Papers with lower-capacity models
One More Modality: Does Abstract Meaning Representation Benefit Visual Question Answering? (2025.findings-emnlp)
Copied to clipboard
| Challenge: | incorporating explicit semantic information, in the form of Abstract Meaning Representation graphs, can enhance VQA models. |
| Approach: | They augment two vision-language models with sentence- and document-level AMRs . they find that in well-resourced settings, models are negatively impacted by AMR . |
| Outcome: | The proposed model improves in well-resourced and low-resource settings with AMR graphs . the model achieves 13.1% relative gain using sentence-level AMRs compared with the smaller model . |